A Systematic Cross-Comparison of Sequence Classifiers

نویسندگان

Binyamin Rosenfeld

Ronen Feldman

Moshe Fresko

چکیده

In the CoNLL 2003 NER shared task, more than two thirds of the submitted systems used a feature-rich representation of the task. Most of them used the maximum entropy principle to combine the features together. Others used large margin linear classifiers, such as SVM and RRM. In this paper, we compare several common classifiers under exactly the same conditions, demonstrating that the ranking of systems in the shared task is due to feature selection and other causes and not due to inherent qualities of the algorithms, which should be ranked otherwise. We demonstrate that whole-sequence models generally outperform local models, and that large margin classifiers generally outperform maximum entropy-based classifiers.

برای دانلود رایگان متن کامل این مقاله و بیش از 32 میلیون مقاله دیگر ابتدا ثبت نام کنید

ثبت نام

اگر عضو سایت هستید لطفا وارد حساب کاربری خود شوید

منابع مشابه

A Preprocessing Technique to Investigate the Stability of Multi-Objective Heuristic Ensemble Classifiers

Background and Objectives: According to the random nature of heuristic algorithms, stability analysis of heuristic ensemble classifiers has particular importance. Methods: The novelty of this paper is using a statistical method consists of Plackett-Burman design, and Taguchi for the first time to specify not only important parameters, but also optimal levels for them. Minitab and Design Expert ...

متن کامل

A comprehensive experimental comparison of the aggregation techniques for face recognition

In face recognition, one of the most important problems to tackle is a large amount of data and the redundancy of information contained in facial images. There are numerous approaches attempting to reduce this redundancy. One of them is information aggregation based on the results of classifiers built on selected facial areas being the most salient regions from the point of view of classificati...

متن کامل

Comparison of MALDI-TOF mass spectrometry data preprocessing techniques and their effect in sample classification

Matrix assisted laser desorption ionization time-of-flight (MALDITOF) is one of the highthroughput mass spectrometry (MS) technologies, producing data requiring an extensive preprocessing before subsequent analyses. In the last years, various preprocessing techniques have been developed for different tasks, including baseline correction, smoothing, normalization, peak detection and peak alignme...

متن کامل

Comparison of Machine Learning Algorithms for Broad Leaf Species Classification Using UAV-RGB Images

Abstract: Knowing the tree species combination of forests provides valuable information for studying the forest’s economic value, fire risk assessment, biodiversity monitoring, and wildlife habitat improvement. Fieldwork is often time-consuming and labor-required, free satellite data are available in coarse resolution and the use of manned aircraft is relatively costly. Recently, unmanned aeria...

متن کامل

دسته‌بندی پرسش‌ها با استفاده از ترکیب دسته‌بندها

Question answering systems are produced and developed to provide exact answers to the question posted in natural language. One of the most important parts of question answering systems is question classification. The purpose of question classification is predicting the kind of answer needed for the question in natural language. The literature works can be categorized as rule-based and learning...

متن کامل

ذخیره در منابع من

ذخیره در منابع من قبلا به منابع من ذحیره شده

{@ msg_add @}

با ذخیره ی این منبع در منابع من، دسترسی به آن را برای استفاده های بعدی آسان تر کنید

عنوان ژورنال:

دوره شماره

صفحات -

تاریخ انتشار 2006

A Systematic Cross-Comparison of Sequence Classifiers

نویسندگان

چکیده

منابع مشابه

A Preprocessing Technique to Investigate the Stability of Multi-Objective Heuristic Ensemble Classifiers

A comprehensive experimental comparison of the aggregation techniques for face recognition

Comparison of MALDI-TOF mass spectrometry data preprocessing techniques and their effect in sample classification

Comparison of Machine Learning Algorithms for Broad Leaf Species Classification Using UAV-RGB Images

دسته‌بندی پرسش‌ها با استفاده از ترکیب دسته‌بندها

عنوان ژورنال:

اشتراک گذاری